Re-Identification Risk versus Data Utility for Aggregated Mobility Research Using Mobile Phone Location Data
نویسندگان
چکیده
Mobile phone location data is a newly emerging data source of great potential to support human mobility research. However, recent studies have indicated that many users can be easily re-identified based on their unique activity patterns. Privacy protection procedures will usually change the original data and cause a loss of data utility for analysis purposes. Therefore, the need for detailed data for activity analysis while avoiding potential privacy risks presents a challenge. The aim of this study is to reveal the re-identification risks from a Chinese city's mobile users and to examine the quantitative relationship between re-identification risk and data utility for an aggregated mobility analysis. The first step is to apply two reported attack models, the top N locations and the spatio-temporal points, to evaluate the re-identification risks in Shenzhen City, a metropolis in China. A spatial generalization approach to protecting privacy is then proposed and implemented, and spatially aggregated analysis is used to assess the loss of data utility after privacy protection. The results demonstrate that the re-identification risks in Shenzhen City are clearly different from those in regions reported in Western countries, which prove the spatial heterogeneity of re-identification risks in mobile phone location data. A uniform mathematical relationship has also been found between re-identification risk (x) and data (y) utility for both attack models: y = -axb+c, (a, b, c>0; 0<x<1), where the exponent b increases with the background knowledge of the attackers. The discovered mathematical relationship provides data publishers with useful guidance on choosing the right tradeoff between privacy and utility. Overall, this study contributes to a better understanding of re-identification risks and a privacy-utility tradeoff benchmark for improving privacy protection when sharing detailed trajectory data.
منابع مشابه
Inferring human mobility using communication patterns
Understanding the patterns of mobility of individuals is crucial for a number of reasons, from city planning to disaster management. There are two common ways of quantifying the amount of travel between locations: by direct observations that often involve privacy issues, e.g., tracking mobile phone locations, or by estimations from models. Typically, such models build on accurate knowledge of t...
متن کاملExtracting Dynamic Urban Mobility Patterns from Mobile Phone Data
The rapid development of information and communication technologies (ICTs) has provided rich resources for spatio-temporal data mining and knowledge discovery in modern societies. Previous research has focused on understanding aggregated urban mobility patterns based on mobile phone datasets, such as extracting activity hotspots and clusters. In this paper, we aim to go one step further from id...
متن کاملUnderstanding Temporal Human Mobility Patterns in a City by Mobile Cellular Data Mining, Case Study: Tehran City
Recent studies have shown that urban complex behaviors like human mobility should be examined by newer and smarter methods. The ubiquitous use of mobile phones and other smart communication devices helps us use a bigger amount of data that can be browsed by the hours of the day, the days of the week, geographic area, meteorological conditions, and so on. In this article, mobile cellular data mi...
متن کاملUnderstanding the Representativeness of Mobile Phone Location Data in Characterizing Human Mobility Indicators
The advent of big data has aided understanding of the driving forces of human mobility, which is beneficial for many fields, such as mobility prediction, urban planning, and traffic management. However, the data sources used in many studies, such as mobile phone location and geo-tagged social media data, are sparsely sampled in the temporal scale. An individual’s records can be distributed over...
متن کاملWi-Fi Mobility Classification on a Mobile Phone for Energy Efficient Activity Tracking
Tracking user location and physical activity is quite common especially among fitness and utility applications on modern smartphones. However, use of GPS and accelerometer sensors to obtain such data is energy consuming and in general cannot be used for extended periods of time. In this paper we describe an approach to detect mobility states and thus turn on and off these energy consuming senso...
متن کامل